Papers by Rana Muhammad Shahroz Khan
The Adaptive Interrogator: Detecting Trojan LLMs in Multi-Agent Systems via Evolved Conversational Strategies (2026.findings-acl)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) have been focused on single-agent, white-box environments, but multi-agend systems (MAS) have a critical blind spot: supply chain vulnerabilities. |
| Approach: | They propose a black-box auditing framework that leverages an Evolutionary Algorithm to autonomously expose hidden threats. |
| Outcome: | The proposed framework achieves superior detection rates (up to 100% in specific configurations) and robustness across diverse architectures. |